Extracting information from S-curves of language change
نویسندگان
چکیده
It is well accepted that adoption of innovations are described by S-curves (slow start, accelerating period and slow end). In this paper, we analyse how much information on the dynamics of innovation spreading can be obtained from a quantitative description of S-curves. We focus on the adoption of linguistic innovations for which detailed databases of written texts from the last 200 years allow for an unprecedented statistical precision. Combining data analysis with simulations of simple models (e.g. the Bass dynamics on complex networks), we identify signatures of endogenous and exogenous factors in the S-curves of adoption. We propose a measure to quantify the strength of these factors and three different methods to estimate it from S-curves. We obtain cases in which the exogenous factors are dominant (in the adoption of German orthographic reforms and of one irregular verb) and cases in which endogenous factors are dominant (in the adoption of conventions for romanization of Russian names and in the regularization of most studied verbs). These results show that the shape of S-curve is not universal and contains information on the adoption mechanism.
منابع مشابه
Presenting a method for extracting structured domain-dependent information from Farsi Web pages
Extracting structured information about entities from web texts is an important task in web mining, natural language processing, and information extraction. Information extraction is useful in many applications including search engines, question-answering systems, recommender systems, machine translation, etc. An information extraction system aims to identify the entities from the text and extr...
متن کاملA Supervised Method for Constructing Sentiment Lexicon in Persian Language
Due to the increasing growth of digital content on the internet and social media, sentiment analysis problem is one of the emerging fields. This problem deals with information extraction and knowledge discovery from textual data using natural language processing has attracted the attention of many researchers. Construction of sentiment lexicon as a valuable language resource is a one of the imp...
متن کاملارائه مدلی برای استخراج اطلاعات از مستندات متنی، مبتنی بر متنکاوی در حوزه یادگیری الکترونیکی
As computer networks become the backbones of science and economy, enormous quantities documents become available. So, for extracting useful information from textual data, text mining techniques have been used. Text Mining has become an important research area that discoveries unknown information, facts or new hypotheses by automatically extracting information from different written documents. T...
متن کاملThe Relationship betweenEFL Learners’ Self-Identity Changes, Motivation Types, and EFL Proficiency
This study aimed to explore the relationships between foreign language learners’ self-identity changes, motivation types, and Foreign Language proficiency associated with learning English in private language schools in Iranian context. Based on a stratified sampling, 204 English as a foreign language learners from three language schools in Tehran were selected to participate in the study. The i...
متن کاملAnalysis of Recorded Rainfall Information for the Purpose of Huff Curves Extraction in the Dez Dam
Identifying the rainfall characteristics and understanding the rainfall-related processes is one of the key factors in the scientific management of water resources. Selection of the design storm is the first step in the estimation of the design flood. Determining temporal rainfall patterns is very important as one of the design rainfall properties in flood estimation and the design of drainage ...
متن کاملذخیره در منابع من
با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید
عنوان ژورنال:
- Journal of the Royal Society, Interface
دوره 11 101 شماره
صفحات -
تاریخ انتشار 2014